Papers with re-identification risks
Anonymity at Risk? Assessing Re-Identification Capabilities of Large Language Models in Court Decisions (2024.findings-naacl)
Copied to clipboard
| Challenge: | Despite high re-identification rates on Wikipedia, even the best LLMs struggled with court decisions. |
| Approach: | They construct an anonymized Wikipedia dataset to investigate re-identification risks . they also introduce new metrics to measure performance . |
| Outcome: | The proposed model can be used to identify individuals in court decisions, but it fails in the vast majority of cases. |